Papers by Mai Hoang Dao
COVID-19 Named Entity Recognition for Vietnamese (2021.naacl-main)
Copied to clipboard
| Challenge: | a new dataset is being developed to help fight the COVID-19 pandemic . the dataset is annotated for the named entity recognition task with newly-defined entity types . |
| Approach: | They present the first manually-annotated COVID-19 domain-specific dataset for Vietnamese . their dataset is annotated for the named entity recognition task with newly-defined entity types . |
| Outcome: | The proposed dataset is the first manually-annotated COVID-19 domain-specific dataset for Vietnamese. |
A Pilot Study of Text-to-SQL Semantic Parsing for Vietnamese (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Semantic parsing is an important NLP task, but Vietnamese is a low-resource language. |
| Approach: | They extend EditSQL and IRNet semantic parsing baselines on Vietnamese datasets . they find automatic Vietnamese word segmentation improves parser results . |
| Outcome: | The proposed dataset improves on two strong parsing baselines for Vietnamese . the monolingual language model PhoBERT improves over the best multilingual language models. |